NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

A Hypothesis Testing-based Cyber Deception Framework under Sequential Reconnaissance

Kumar, Abhinav; Brahma, Swastik; Kamhoua, Charles A (March 2025, Annual Conference on Information Sciences and Systems (CISS))

Free, publicly-accessible full text available March 19, 2026
Adaptive Incentive Design for Markov Decision Processes with Unknown Rewards

Ma, Haoxiang; Han, Shuo; Hemida, Ahmed; Kamhoua, Charles A; Fu, Jie (March 2025, Transactions on Machine Learning Research)
Poupart, Pascal (Ed.)
Incentive design, also known as model design or environment design for Markov decision processes(MDPs), refers to a class of problems in which a leader can incentivize his follower by modifying the follower's reward function, in anticipation that the follower's optimal policy in the resulting MDP can be desirable for the leader's objective. In this work, we propose gradient-ascent algorithms to compute the leader's optimal incentive design, despite the lack of knowledge about the follower's reward function. First, we formulate the incentive design problem as a bi-level optimization problem and demonstrate that, by the softmax temporal consistency between the follower's policy and value function, the bi-level optimization problem can be reduced to single-level optimization, for which a gradient-based algorithm can be developed to optimize the leader's objective. We establish several key properties of incentive design in MDPs and prove the convergence of the proposed gradient-based method. Next, we show that the gradient terms can be estimated from observations of the follower's best response policy, enabling the use of a stochastic gradient-ascent algorithm to compute a locally optimal incentive design without knowing or learning the follower's reward function. Finally, we analyze the conditions under which an incentive design remains optimal for two different rewards which are policy invariant. The effectiveness of the proposed algorithm is demonstrated using a small probabilistic transition system and a stochastic gridworld.
more » « less
Free, publicly-accessible full text available March 28, 2026
A Hypothesis Testing-based Framework for Cyber Deception with Sludging

https://doi.org/10.1109/Allerton63246.2024.10735267

Kumar, Abhinav; Brahma, Swastik; Geng, Baocheng; Kamhoua, Charles A; Varshney, Pramod K (September 2024, 2024 60th Annual Allerton Conference on Communication, Control, and Computing)

Full Text Available
Optimizing Sensor Allocation Against Attackers With Uncertain Intentions: A Worst-Case Regret Minimization Approach

https://doi.org/10.1109/LCSYS.2023.3290489

Ma, Haoxiang; Han, Shuo; Kamhoua, Charles A.; Fu, Jie (January 2023, IEEE Control Systems Letters)

This letter focuses on the optimal allocation of multi-stage attacks with the uncertainty in attacker’s intention. We model the attack planning problem using a Markov decision process and characterize the uncertainty in the attacker’s intention using a finite set of reward functions—each reward represents a type of attacker. Based on this modeling, we employ the paradigm of the worst-case absolute regret minimization from robust game theory and develop mixed-integer linear program (MILP) formulations for solving the worst-case regret minimizing sensor allocation strategies for two classes of attack-defend interactions: one where the defender and attacker engage in a zero-sum game and another where they engage in a non-zero-sum game. We demonstrate the effectiveness of our algorithm using a stochastic gridworld example.
more » « less
Full Text Available
Game and Prospect Theoretic Hardware Trojan Testing

Nan, Satyaki; Njilla, Laurent; Brahma, Swastik; Kamhoua, Charles A. (January 2023, Annual Conference on Information Sciences and Systems)

Full Text Available
Synthesizing Attack-Aware Control and Active Sensing Strategies Under Reactive Sensor Attacks

https://doi.org/10.1109/LCSYS.2022.3187313

Udupa, Sumukha; Kulkarni, Abhishek N.; Han, Shuo; Leslie, Nandi O.; Kamhoua, Charles A.; Fu, Jie (January 2023, IEEE Control Systems Letters)

Full Text Available
Mitigation of Jamming Attacks via Deception

https://doi.org/10.1109/PIMRC48278.2020.9217140

Nan, Satyaki; Brahma, Swastik; Kamhoua, Charles A.; Leslie, Nandi O. (October 2020, IEEE International Symposium on Personal, Indoor and Mobile Radio Communications)
null (Ed.)
Full Text Available

Search for: All records